Papers with diagnostic reasoning task

1 papers
Experiments or Outcomes? Probing Scientific Feasibility in Large Language Models (2026.acl-short)

Copied to clipboard

Challenge: Scientific feasibility assessment asks whether a claim aligns with established knowledge and whether experimental evidence could support or refute it.
Approach: They frame scientific feasibility assessment as a diagnostic reasoning task . given a hypothesis, a model predicts feasible or infeasible and justifies its decision . they evaluate large language models under controlled knowledge conditions .
Outcome: The results show that providing outcome evidence is more reliable than providing experiment descriptions.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations